T-Cube: Quick Response to Ad-Hoc Time Series Queries against Large Datasets
نویسندگان
چکیده
We present a novel data structure called T-Cube which dramatically improves response time to ad-hoc time series queries against large datasets. We have tested T-Cube on both synthetic and real world data (emergency room patient visits, pharmacy sales) containing millions of records. The results indicate that T-Cube responds to complex queries 1,000 times faster when compared to the state-of-the-art commercial time series extraction tools. This speedup has two main benefits: (1) It enables massive scale statistical data mining of large collections of time series, and (2) It allows its users to perform many complex ad-hoc queries without inconvenient delays.
منابع مشابه
T-Cube: Data Structure for Fast Extraction of Time Series from Large Datasets
This report introduces a data structure called T-Cube designed to dramatically improve response time to ad-hoc time series queries against large datasets. We have tested T-Cube on both synthetic and real world data (emergency room patient visits, pharmacy sales) containing millions of records. The results indicate that T-Cube responds to complex queries 1,000 times faster when compared to the s...
متن کاملارائه روشی پویا جهت پاسخ به پرسوجوهای پیوسته تجمّعی اقتضایی
Data Streams are infinite, fast, time-stamp data elements which are received explosively. Generally, these elements need to be processed in an online, real-time way. So, algorithms to process data streams and answer queries on these streams are mostly one-pass. The execution of such algorithms has some challenges such as memory limitation, scheduling, and accuracy of answers. They will be more ...
متن کاملData Cubes in Dynamic Environments
The data cube, also known in the OLAP community as the multidimensional database, is designed to provide aggregate information that can be used to analyze the contents of databases and data warehouses. Previous research mainly focussed on strategies for supporting queries, assuming that updates do not play an important role and can be propagated to the data cube in batches. While this might be ...
متن کاملKnowledge and Information Systems REGULAR PAPER
In some business applications such as trading management in financial institutions, it is required to accurately answer ad hoc aggregate queries over data streams. Materializing and incrementally maintaining a full data cube or even its compression or approximation over a data stream is often computationally prohibitive. On the other hand, although previous studies proposed approximate methods ...
متن کاملA Recursive Relative Prefix Sum Approach to Range Sum Queries in Data Warehouses
Data warehouses contain data consolidated from several operational databases and provide the historical, and summarized data. On-Line Analytical Processing (OLAP) is designed to provide aggregate information to analyze the contents of data warehouses. An increasingly popular data model for OLAP applications is the multidimensional database, also known as data cube. A range sum query applies a s...
متن کاملذخیره در منابع من
با ذخیره ی این منبع در منابع من، دسترسی به آن را برای استفاده های بعدی آسان تر کنید
عنوان ژورنال:
دوره شماره
صفحات -
تاریخ انتشار 2007